首页> 外文OA文献 >MetaLDA: a Topic Model that Efficiently Incorporates Meta information

【2h】

MetaLDA: a Topic Model that Efficiently Incorporates Meta information

机译：metaLDa：一种有效整合元信息的主题模型

代理获取

本网站仅为用户提供外文OA文献查询和代理获取服务，本网站没有原文。下单后我们将采用程序或人工为您竭诚获取高质量的原文，但由于OA文献来源多样且变更频繁，仍可能出现获取不到、文献不完整或与标题不符等情况，如果获取不到我们将提供退款服务。请知悉。

页面导航

摘要
著录项
相似文献
相关主题

摘要

Besides the text content, documents and their associated words usually comewith rich sets of meta informa- tion, such as categories of documents andsemantic/syntactic features of words, like those encoded in word embeddings.Incorporating such meta information directly into the generative process oftopic models can improve modelling accuracy and topic quality, especially inthe case where the word-occurrence information in the training data isinsufficient. In this paper, we present a topic model, called MetaLDA, which isable to leverage either document or word meta information, or both of themjointly. With two data argumentation techniques, we can derive an efficientGibbs sampling algorithm, which benefits from the fully local conjugacy of themodel. Moreover, the algorithm is favoured by the sparsity of the metainformation. Extensive experiments on several real world datasets demonstratethat our model achieves comparable or improved performance in terms of bothperplexity and topic quality, particularly in handling sparse texts. Inaddition, compared with other models using meta information, our model runssignificantly faster.

机译：除了文本内容之外，文档及其相关单词通常还带有丰富的元信息集，例如文档类别和单词的语义/句法特征，例如在单词嵌入中编码的那些。将此类元信息直接纳入主题模型的生成过程中可以提高建模准确性和主题质量，尤其是在训练数据中的单词出现信息不足的情况下。在本文中，我们提出了一个称为MetaLDA的主题模型，该模型能够利用文档或单词元信息，或者同时使用这两者。利用两种数据论证技术，我们可以得出一种有效的Gibbs采样算法，该算法得益于模型的完全局部共轭性。而且，该算法因元信息的稀疏性而受到青睐。在几个真实世界的数据集上进行的大量实验表明，我们的模型在困惑度和主题质量方面都达到了可比或改善的性能，尤其是在处理稀疏文本方面。此外，与使用元信息的其他模型相比，我们的模型运行速度显着提高。

著录项

作者
Zhao, He; Du, Lan; Buntine, Wray; Liu, Gang;
展开▼
作者单位

展开▼
年度 2017
总页数
原文格式 PDF
正文语种
中图分类

相似文献

外文文献
中文文献
专利

1. Efficient Correlated Topic Modeling with Topic Embedding [J] . Junxian He, Zhiting Hu, Taylor Berg-Kirkpatrick, SIGKDD explorations . 2017,第CDaROM期

机译：主题嵌入有效相关主题建模
2. Enhancing Topic Models by Incorporating Explicit and Implicit External Knowledge [J] . Yang Hong, Xinhuai Tang, Tiancheng Tang, JMLR: Workshop and Conference Proceedings . 2020,第2010期

机译：通过结合显式和隐式外部知识来增强主题模型
3. Incorporating word embeddings into topic modeling of short text [J] . Gao Wang, Peng Min, Wang Hua, Knowledge and information systems . 2019,第2期

机译：将Word Embeddings纳入了短文本的主题建模
4. MetaLDA: A Topic Model that Efficiently Incorporates Meta Information [C] . He Zhao, Lan Du, Wray Buntine, IEEE International Conference on Data Mining . 2017

机译：MetaLDA：有效整合元信息的主题模型
5. Three Research Topics in Education: (1) Associations between Approaches to Learning and Academic Achievement; (2) a Meta- Analytic Review on Approaches to Learning and Academic Achievement; (3) Power Analysis in Meta-Analysis: A Three-Level Model [D] . Zhang, Bixi. 2021

机译：教育三大研究主题：（1）学习途径与学术成果之间的协会; （2）关于学习和学术成就的方法的元分析审查; （3）Meta分析中的功率分析：三级模型
6. Topic Modeling for Analyzing Patients’ Perceptions and Concerns of Hearing Loss on Social QA Sites: Incorporating Patients’ Perspective [O] . Junghwa Bahng, Chang Heon Lee 2020

机译：分析患者对社会问答障碍损失患者的看法和担忧的主题建模：纳入患者的观点
7. Efficient Methods for Incorporating Knowledge into Topic Models [O] . Yi Yang, Doug Downey, Jordan Boyd-graber 2015

机译：将知识纳入主题模型的有效方法

MetaLDA: a Topic Model that Efficiently Incorporates Meta information

摘要

著录项

相似文献

相关主题

期刊订阅